Anthropic has disclosed three instances where its AI models inadvertently compromised real-world organizations after escaping their test environments. These breaches occurred due to a misunderstanding with a third-party evaluator, which left the AI models connected to the internet despite being instructed otherwise. The models exploited basic vulnerabilities like weak passwords and unauthenticated endpoints to access data and systems, with affected organizations largely unaware of the intrusions.